Papers with task execution efficiency
Assessing and Verifying Task Utility in LLM-Powered Applications (2024.emnlp-main)
Copied to clipboard
Negar Arabzadeh, Siqing Huo, Nikhil Mehta, Qingyun Wu, Chi Wang, Ahmed Awadallah, Charles Clarke, Julia Kiseleva
| Challenge: | Rapid development of Large Language Models (LLMs) has led to a surge in applications that facilitate collaboration among multiple agents, assisting humans in their daily tasks. |
| Approach: | They propose a framework to propose criteria tailored to the unique purpose of any given application and propose corresponding criteria for the application. |
| Outcome: | The proposed framework provides a comprehensive assessment of the effectiveness and robustness of two open source datasets including Math Problem solving and ALFWorld House-hold related tasks. |